About me


Senior Researcher

I perform my main activity as a Senior researcher at Sandvik RTD (Tampere, Finland), while maintaining collaboration and relationship with the Machine Listening Group at Tampere Universiy.

My research interests include audio signal processing, machine learning, acoustic scene classification, sound event detection and audio captioning.

My career

My research trajectory as Postdoctoral researcher began in the Image and Signal Processing group. In 2020, I move to Tampere, serving for several years as a full-time Postdoctoral Research Fellow alongside Assistant professor Annamaria Mesaros at the Machine Listening Group. I received my PhD in 2019 from the Universitat de València, where I worked in the signal processing and acoustic technology group. I have worked on acoustic signal processing, more specifically in the field of sound event recognition, the topic of my PhD thesis.

I have diverse research interests and I am always happy to dive into a new area.

Datasets

  • AVCaps: An audio-visual dataset with modality-specific captions, contains modality specific captions from 2061 video clips spanning a total of 28.8 hours.
  • MAESTRO Real: Multiple Annotator Estimated STROng labels, contains 49 audio files from 5 different real acoustic scenes, each of them from 3 to 5 min long.
  • MAESTRO Synthetic: Multiple Annotator Estimated STROng labels, contains 20 audio files created using Scaper, each of them 3 min long.
  • MACS: Multi-Annotator Captioned Soundscapes, contains audio captions and corresponding audio tags for a number of 3930 audio files of the TAU Urban Acoustic Scenes 2019 development dataset (airport, public square and park)
  • MATS: Multi-Annotator Tagged Soundscapes, contains audio tags from 133 annotators for a number of 3930 audio files of the TAU Urban Acoustic Scenes 2019 development dataset (airport, public square and park)